NAP: Practical Fault-Tolerance for Itinerant Computations

نویسندگان

  • Dag Johansen
  • Keith Marzullo
  • Fred B. Schneider
  • Kjetil Jacobsen
  • Dmitrii Zagorodnov
چکیده

NAP, a detection and recovery based scheme for implementing fault-tolerant itinerant computations, is presented. We give the semantics for the scheme and describe a protocol that implements NAP in tacoma.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

An approach to fault detection and correction in design of systems using of Turbo ‎codes‎

We present an approach to design of fault tolerant computing systems. In this paper, a technique is employed that enable the combination of several codes, in order to obtain flexibility in the design of error correcting codes. Code combining techniques are very effective, which one of these codes are turbo codes. The Algorithm-based fault tolerance techniques that to detect errors rely on the c...

متن کامل

Exploiting Soft Computing for Increased Fault Tolerance

Traditionally, fault tolerance researchers have made very strict assumptions about program correctness. Such strict notions of correctness are appropriate for workloads that are numerically oriented. However, a growing number of important workloads produce results that have a higher (often qualitative) user-level interpretation. We call these soft computations. Examples of soft computations inc...

متن کامل

Runtime system level fault tolerance for a distributed functional language

Distributed Fault Tolerance entails detecting errors, confining the damage caused, recovery from the errors, and providing continued service on a network of co-operating machines. Functional languages potentially offer benefits for distributed fault tolerance: many computations are pure, and hence have no side-effects to be reversed during error recovery. Moreover functional languages have a hi...

متن کامل

Fault Tolerant Computation of Large Inner Products Fault Tolerant Computation of Large Inner Products

In this paper we introduce a new technique for applying fault tolerance to Modulus Replication RNS computations by adding redundancy to the independent computational channels. This technique provides a low-overhead solution to fault tolerant large inner product computations.

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 1999